NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

CleAR: Robust Context-Guided Generative Lighting Estimation for Mobile Augmented Reality

https://doi.org/10.1145/3749535

Zhao, Yiqin; Dasari, Mallesham; Guo, Tian (September 2025, Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies)

High-quality environment lighting is essential for creating immersive mobile augmented reality (AR) experiences. However, achieving visually coherent estimation for mobile AR is challenging due to several key limitations in AR device sensing capabilities, including low camera FoV and limited pixel dynamic ranges. Recent advancements in generative AI, which can generate high-quality images from different types of prompts, including texts and images, present a potential solution for high-quality lighting estimation. Still, to effectively use generative image diffusion models, we must address two key limitations of content quality and slow inference. In this work, we design and implement a generative lighting estimation system called CleAR that can produce high-quality, diverse environment maps in the format of 360° HDR images. Specifically, we design a two-step generation pipeline guided by AR environment context data to ensure the output aligns with the physical environment's visual context and color appearance. To improve the estimation robustness under different lighting conditions, we design a real-time refinement component to adjust lighting estimation results on AR devices. To train and test our generative models, we curate a large-scale environment lighting estimation dataset with diverse lighting conditions. Through a combination of quantitative and qualitative evaluations, we show that CleAR outperforms state-of-the-art lighting estimation methods on both estimation accuracy, latency, and robustness, and is rated by 31 participants as producing better renderings for most virtual objects. For example, CleAR achieves 51% to 56% accuracy improvement on virtual object renderings across objects of three distinctive types of materials and reflective properties. CleAR produces lighting estimates of comparable or better quality in just 3.2 seconds---over 110X faster than state-of-the-art methods. Moreover, CleAR supports real-time refinement of lighting estimation results, ensuring robust and timely updates for AR applications.
more » « less
Free, publicly-accessible full text available September 3, 2026
Towards In-context Environment Sensing for Mobile Augmented Reality

https://doi.org/10.1145/3636534.3696211

Zhao, Yiqin; Ganj, Ashkan; Guo, Tian (December 2024, ACM)

Full Text Available
Demo: ARFlow: A Framework for Simplifying AR Experimentation Workflow

https://doi.org/10.1145/3638550.3643617

Zhao, Yiqin; Guo, Tian (February 2024, ACM)
Mobile AR Depth Estimation: Challenges & Prospects

https://doi.org/10.1145/3638550.3641122

Ganj, Ashkan; Zhao, Yiqin; Su, Hang; Guo, Tian (February 2024, ACM)
Toward Scalable and Controllable AR Experimentation

https://doi.org/10.1145/3615452.3617941

Ganj, Ashkan; Zhao, Yiqin; Galbiati, Federico; Guo, Tian (October 2023, ACM)
Multi-Camera Lighting Estimation for Photorealistic Front-Facing Mobile Augmented Reality

https://doi.org/10.1145/3572864.3580337

Zhao, Yiqin; Fanello, Sean; Guo, Tian (February 2023, HotMobile '23: Proceedings of the 24th International Workshop on Mobile Computing Systems and Applications)

Lighting understanding plays an important role in virtual object composition, including mobile augmented reality (AR) applications. Prior work often targets recovering lighting from the physical environment to support photorealistic AR rendering. Because the common workflow is to use a back-facing camera to capture the physical world for overlaying virtual objects, we refer to this usage pattern as back-facing AR. However, existing methods often fall short in supporting emerging front-facing mobile AR applications, e.g., virtual try-on where a user leverages a front-facing camera to explore the effect of various products (e.g., glasses or hats) of different styles. This lack of support can be attributed to the unique challenges of obtaining 360° HDR environment maps, an ideal format of lighting representation, from the front-facing camera and existing techniques. In this paper, we propose to leverage dual-camera streaming to generate a high-quality environment map by combining multi-view lighting reconstruction and parametric directional lighting estimation. Our preliminary results show improved rendering quality using a dual-camera setup for front-facing AR compared to a commercial solution.
more » « less
Full Text Available
Privacy-preserving Reflection Rendering for Augmented Reality

https://doi.org/10.1145/3503161.3548386

Zhao, Yiqin; Wei, Sheng; Guo, Tian (October 2022, ACM International Conference on Multimedia)

Full Text Available
LITAR: Visually Coherent Lighting for Mobile Augmented Reality

https://doi.org/10.1145/3550291

Zhao, Yiqin; Ma, Chongyang; Huang, Haibin; Guo, Tian (September 2022, Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies)

An accurate understanding of omnidirectional environment lighting is crucial for high-quality virtual object rendering in mobile augmented reality (AR). In particular, to support reflective rendering, existing methods have leveraged deep learning models to estimate or have used physical light probes to capture physical lighting, typically represented in the form of an environment map. However, these methods often fail to provide visually coherent details or require additional setups. For example, the commercial framework ARKit uses a convolutional neural network that can generate realistic environment maps; however the corresponding reflective rendering might not match the physical environments. In this work, we present the design and implementation of a lighting reconstruction framework called LITAR that enables realistic and visually-coherent rendering. LITAR addresses several challenges of supporting lighting information for mobile AR. First, to address the spatial variance problem, LITAR uses two-field lighting reconstruction to divide the lighting reconstruction task into the spatial variance-aware near-field reconstruction and the directional-aware far-field reconstruction. The corresponding environment map allows reflective rendering with correct color tones. Second, LITAR uses two noise-tolerant data capturing policies to ensure data quality, namely guided bootstrapped movement and motion-based automatic capturing. Third, to handle the mismatch between the mobile computation capability and the high computation requirement of lighting reconstruction, LITAR employs two novel real-time environment map rendering techniques called multi-resolution projection and anchor extrapolation. These two techniques effectively remove the need of time-consuming mesh reconstruction while maintaining visual quality. Lastly, LITAR provides several knobs to facilitate mobile AR application developers making quality and performance trade-offs in lighting reconstruction. We evaluated the performance of LITAR using a small-scale testbed experiment and a controlled simulation. Our testbed-based evaluation shows that LITAR achieves more visually coherent rendering effects than ARKit. Our design of multi-resolution projection significantly reduces the time of point cloud projection from about 3 seconds to 14.6 milliseconds. Our simulation shows that LITAR, on average, achieves up to 44.1% higher PSNR value than a recent work Xihe on two complex objects with physically-based materials.
more » « less
Full Text Available
FusedAR: Adaptive Environment Lighting Reconstruction for Visually Coherent Mobile AR Rendering

https://doi.org/10.1109/VRW55335.2022.00137

Zhao, Yiqin; Guo, Tian (January 2022, IEEE Conference on Virtual Reality and 3D User Interfaces Abstracts and Workshops (VRW) (IEEEVRW'22))

Full Text Available
Xihe: a 3D vision-based lighting estimation framework for mobile augmented reality

https://doi.org/10.1145/3458864.3467886

Zhao, Yiqin; Guo, Tian (June 2021, MobiSys '21: Proceedings of the 19th Annual International Conference on Mobile Systems, Applications, and Services)
null (Ed.)
Full Text Available

« Prev Next »

Search for: All records